motion space
Hamiltonian latent operators for content and motion disentanglement in image sequences
We introduce HALO - a deep generative model utilising HAmiltonian Latent Operators to reliably disentangle content and motion information in image sequences. The content represents summary statistics of a sequence, and motion is a dynamic process that determines how information is expressed in any part of the sequence. By modelling the dynamics as a Hamiltonian motion, important desiderata are ensured: (1) the motion is reversible, (2) the symplectic, volume-preserving structure in phase space means paths are continuous and are not divergent in the latent space. Consequently, the nearness of sequence frames is realised by the nearness of their coordinates in the phase space, which proves valuable for disentanglement and long-term sequence generation. The sequence space is generally comprised of different types of dynamical motions. To ensure long-term separability and allow controlled generation, we associate every motion with a unique Hamiltonian that acts in its respective subspace. We demonstrate the utility of HALO by swapping the motion of a pair of sequences, controlled generation, and image rotations.
Driving Animatronic Robot Facial Expression From Speech
Li, Boren, Li, Hang, Liu, Hangxin
Animatronic robots aim to enable natural human-robot interaction through lifelike facial expressions. However, generating realistic, speech-synchronized robot expressions is challenging due to the complexities of facial biomechanics and responsive motion synthesis. This paper presents a principled, skinning-centric approach to drive animatronic robot facial expressions from speech. The proposed approach employs linear blend skinning (LBS) as the core representation to guide tightly integrated innovations in embodiment design and motion synthesis. LBS informs the actuation topology, enables human expression retargeting, and allows speech-driven facial motion generation. The proposed approach is capable of generating highly realistic, real-time facial expressions from speech on an animatronic face, significantly advancing robots' ability to replicate nuanced human expressions for natural interaction.
Sandwich Approach for Motion Planning and Control
Ramezani, Mohamadreza, Rastgoftar, Hossein
This paper develops a new approach for robot motion planning and control in obstacle-laden environments that is inspired by fundamentals of fluid mechanics. For motion planning, we propose a novel transformation between motion space, with arbitrary obstacles of random sizes and shapes, and an obstacle-free planning space with geodesically-varying distances and constrained transitions. We then obtain robot desired trajectory by A* searching over a uniform grid distributed over the planning space. We show that implementing the A* search over the planning space can generate shorter paths when compared to the existing A* searching over the motion space. For trajectory tracking, we propose an MPC-based trajectory tracking control, with linear equality and inequality safety constraints, enforcing the safety requirements of planning and control.
Classifying Patterns of Visual Motion - a Neuromorphic Approach
Heinzle, Jakob, Stocker, Alan A.
We report a system that classifies and can learn to classify patterns of visual motion online. The complete system is described by the dynamics of its physical network architectures. The combination of the following properties makes the system novel: Firstly, the front-end of the system consists of an aVLSI optical flow chip that collectively computes 2-D global visual motion in real-time [1]. Secondly, the complexity of the classification task is significantly reduced by mapping the continuous motion trajectories to sequences of'motion events'. And thirdly, all the network structures are simple and with the exception of the optical flow chip based on a Winner-Take-All (WTA) architecture. We demonstrate the application of the proposed generic system for a contactless man-machine interface that allows to write letters by visual motion. Regarding the low complexity of the system, its robustness and the already existing front-end, a complete aVLSI system-on-chip implementation is realistic, allowing various applications in mobile electronic devices.
Classifying Patterns of Visual Motion - a Neuromorphic Approach
Heinzle, Jakob, Stocker, Alan A.
We report a system that classifies and can learn to classify patterns of visual motion online. The complete system is described by the dynamics of its physical network architectures. The combination of the following properties makes the system novel: Firstly, the front-end of the system consists of an aVLSI optical flow chip that collectively computes 2-D global visual motion in real-time [1]. Secondly, the complexity of the classification task is significantly reduced by mapping the continuous motion trajectories to sequences of'motion events'. And thirdly, all the network structures are simple and with the exception of the optical flow chip based on a Winner-Take-All (WTA) architecture. We demonstrate the application of the proposed generic system for a contactless man-machine interface that allows to write letters by visual motion. Regarding the low complexity of the system, its robustness and the already existing front-end, a complete aVLSI system-on-chip implementation is realistic, allowing various applications in mobile electronic devices.
Classifying Patterns of Visual Motion - a Neuromorphic Approach
Heinzle, Jakob, Stocker, Alan A.
We report a system that classifies and can learn to classify patterns of visual motion online. The complete system is described by the dynamics ofits physical network architectures. The combination of the following propertiesmakes the system novel: Firstly, the front-end of the system consists of an aVLSI optical flow chip that collectively computes 2-D global visual motion in real-time [1]. Secondly, the complexity of the classification task is significantly reduced by mapping the continuous motiontrajectories to sequences of'motion events'. And thirdly, all the network structures are simple and with the exception of the optical flow chip based on a Winner-Take-All (WTA) architecture. We demonstrate theapplication of the proposed generic system for a contactless man-machine interface that allows to write letters by visual motion. Regarding thelow complexity of the system, its robustness and the already existing front-end, a complete aVLSI system-on-chip implementation is realistic, allowing various applications in mobile electronic devices.